European Journal of Nuclear Medicine and Molecular Imaging
○ Springer Science and Business Media LLC
Preprints posted in the last 7 days, ranked by how well they match European Journal of Nuclear Medicine and Molecular Imaging's content profile, based on 20 papers previously published here. The average preprint has a 0.02% match score for this journal, so anything above that is already an above-average fit.
Liu, Z.; Zhao, C.; Huang, Z.; Guo, F.; Wang, D. J.; Shao, X.
Show abstract
Purpose: To develop an accelerated motion-compensated diffusion-weighted pseudo-continuous arterial spin labeling (MCDW-pCASL) method using a spatial subspace low-rank reconstruction method for efficient quantification of blood-brain barrier (BBB) water exchange (kw) and permeability (PSw). Methods: An accelerated multidelay MCDW-pCASL sequence was developed to simultaneously encode intravascular and extravascular diffusion-weighted ASL signals across multiple post-labeling delays (PLDs). A spatial subspace low-rank reconstruction framework was optimized to enable joint estimation of cerebral blood flow (CBF) and BBB water exchange rate and permeability. Fourteen young healthy adults underwent test-retest scans (separated by ~1 week) at 3T with both the accelerated MCDW-pCASL and a conventional diffusion-prepared (DP) pCASL sequence. Whole-brain, gray-matter, and white-matter CBF and kw values were quantified to assess test-retest repeatability and cross-method agreement. An additional cohort of 30 older adults underwent single-session MCDW and DP scans to evaluate age-related perfusion and BBB kw/PSw differences. Intraclass correlation coefficients (ICCs) were used to assess reliability and agreement. Results: Accelerated MCDW-pCASL demonstrated excellent agreement with DP-pCASL for CBF (ICC = 0.89) and fair agreement for kw (ICC = 0.56). Test-retest repeatability of MCDW-pCASL was good for CBF, BBB kw and PSw (ICC {approx} 0.6). Across both sequences, younger subjects exhibited significantly higher CBF and kw compared with older adults. Conclusion: Incorporating a spatial low-rank subspace reconstruction enables accelerated MCDW-pCASL acquisition with reliable simultaneous quantification of CBF, BBB kw and PSw. Clinical applications of this method for assessing perfusion and BBB function are warranted.
d'Angremont, E.; Marschall, T. M.; Renken, R. J.; Sommer, I. E.
Show abstract
Introduction Parkinson's disease (PD) is a multifactorial disorder, affecting multiple neurotransmitter systems, including the cholinergic system. Cholinergic denervation is heterogeneous across patients and difficult to predict based on clinical presentation. In this study, we assessed the sensitivity of structural MRI (sMRI) and functional MRI (fMRI) to cholinergic degeneration related to PD and to cognitive functioning in PD. We compared our results to results from previously reported [18F]Fluoroethoxybenzovesamicol ([18F]FEOBV) PET imaging, which is considered the gold standard for cholinergic imaging. Methods 34 PD patients and 10 healthy controls underwent structural T1-weighted MRI. A subset of 14 patients and 9 controls also underwent resting-state fMRI. We extracted the bilateral volumes of the nucleus basalis of Meynert (NBM) from the sMRI images. Functional connectivity (FC) from the NBM to the cortex (NBM-FC) was determined using fMRI data. Principal component analysis (PCA) was applied to reduce the dimensionality of the NBM-FC images. We assessed performances for NBM-FC in distinguishing patients from controls using stepwise logistic regression. Similarly, NBM volume was used using logistic regression. Furthermore, the relation between these measures and cognitive function in several domains was investigated with (stepwise) linear regression. Leave-one-out cross validation (LOOCV) and bootstrapping was performed to assess robustness of the results. Results NBM-FC was well able to discriminate patients from controls with an AUC of 0.84 (95% CI: 0.62-1). NBM volume showed lower performance, but was still better than chance: AUC: 0.75 (95% CI: 0.57-0.93). Significant correlations were found between 1) cognition in the attentional domain and NBM-FC (r=0.63; p=.015) and 2) global cognition and NBM volume (r=0.55, p=.001). These results were inferior to those previously reported using [18F]FEOBV tracer uptake (see Chapter 6). Bootstrapping revealed that NBM volume of only the left hemisphere was stably related to PD diagnosis and global cognition in PD patients. We found that a lower NBM-FC in specific brain areas, including the fusiform gyrus, supramarginal gyrus and dorsolateral prefrontal cortex, was related to PD diagnosis. Bootstrapping revealed no stable NBM-FC pattern related to attention. Conclusion Although MRI results were slightly inferior to [18F]FEOBV PET data, MRI may provide a cheaper and more widely available alternative for cholinergic imaging. We recommend testing the utility of MRI as predictor and monitor of cholinergic treatment effect in a longitudinal study.
Mavromati, K.; Dyer, A. H.; Beazer, J. D.; Hughes, L.; Kennelly, S. P.; Quinn, T. J.
Show abstract
Background: Plasma phosphorylated tau-217 (pTau-217) measurements for use in Alzheimer disease (AD) identification require thresholds to define positivity and there exist different approaches to operationally defining the boundary. We compared amyloid {beta} (AB) PET-anchored and distribution-based positivity cut-off values and explored how these mapped onto latent biomarker states. Methods: We analysed plasma pTau-217 measured in the Bio-Hermes-001 cohort (N = 990) using an immunoassay (Lilly) and mass spectrometry assay (University of Gothenburg). Gaussian mixture models were used to identify latent classes and thresholds were derived in two ways: achieving 90% specificity for AB PET positivity and exceeding the mean + 2SDs of the lowest latent class. We explore classes in reference to AB PET status and clinical diagnosis, as well as agreement between approaches using Cohen kappa for both assays. Results: In both assays, three latent biomarker classes were identified with monotonic increases in AD clinical diagnosis and AB PET positivity. PET-anchored thresholds showed lower specificity but higher sensitivity to amyloid positivity than distribution-based thresholds. Overall agreement between the approaches was acceptable (k = 0.678 for Lilly and 0.575 for University of Gothenburg), with disagreement concentrated in the intermediate latent class. Classes with the lowest and highest pTau-217 concentrations were classified consistently using both thresholds Discussion: The two thresholding approaches yielded similar classifications at both the negative and positive tail of the observed biomarker distribution, but classify intermediate concentrations differently. The boundary definition influenced pTau-217 positivity more than the analytical platform itself. Thresholding approaches may capture different pTau-217 biomarker states, therefore such methodological decisions should be grounded in the context of the intended application.
Tzanis, E.; Klontzas, M. E.
Show abstract
This study presents ReCo (Research Cosmos), a self-configuring and self-extending agentic research framework for the biomedical domain. ReCo is orchestrated by a large language model that interacts with native computing tools, bundled Model Context Protocol (MCP) servers, structured skills, persistent project memory, and a desktop interface. Its bundled MCP servers provide biomedical analysis capabilities while serving as implementation paradigms for integrating new computational and AI frameworks. Structured skills encode procedures for environment configuration and framework ingestion, enabling ReCo to inspect repositories, manuscripts, or local codebases; identify dependencies and execution patterns; create isolated runtime environments; design and implement MCP interfaces. Self-extension was evaluated using five heterogeneous systems: the Merlin computed tomography foundation model, MAISI-v2 medical image synthesis framework, asari liquid chromatography-mass spectrometry workflow, DosimeTron agentic radiation-dosimetry platform, and Orthanc DICOM server. ReCo successfully operationalized all five systems and completed predefined functional evaluations. Re-hosted DosimeTron outputs demonstrated near-perfect agreement with the reference pipeline across 651 organ observations (Pearson correlation and Lin concordance correlation coefficient, 0.99999; mean absolute percentage difference, 0.37%). Notably, ReCo configured Orthanc as a PACS-like coordination layer, integrated it with DosimeTron, Merlin, and TotalSegmentator, and orchestrated data retrieval, analysis, and return of valid DICOM RTSTRUCT, RTDOSE, and Structured Report. ReCo provides a unified environment for configuring, documenting, and operationalizing heterogeneous biomedical frameworks, reducing technical barriers to the adoption and integration of emerging computational and AI methods. The official open-source ReCo GitHub repository is available at: https://github.com/eltzanis/ReCo
Thommana, A. A.; Donnay, C. A.; Norato, G.; Gaitan, M. I.; Griffanti, L.; Nair, G.; Reich, D. S.; Okar, S. V.
Show abstract
White matter lesion (WML) identification, assessment, and characterization using magnetic resonance imaging (MRI) are fundamental for diagnosis and monitoring of multiple sclerosis (MS). Portable ultra-low field (pULF) MRI at 64 millitesla (mT) has been shown to visualize WML with at least one dimension greater than 4 mm. An automated WML segmentation tool catered to pULF-MRI can provide standardized and accurate quantitative measurements of WML volume. In this study, we sought to investigate and compare the accuracy of machine-learning (ML) and deep-learning (DL) pULF MRI segmentation tools. Same-day paired pULF (64mT) and high-field (HF, 3T) MRI scans from 84 adults with MS or suspected-MS (mean age {+/-} SD: 48 {+/-} 13, 62 females) included T2-FLAIR and T1w images. Reference WML segmentations were manually annotated on pULF T2-FLAIR for all scans, with WML confirmed with registered HF T2-FLAIR. HF reference WML segmentations were created. Four automated segmentation methods were applied to pULF scans: Method for Inter-Modal Segmentation Analysis (MIMoSA), an ML algorithm trained on HF WML masks; WMH-SynthSeg, a convolutional neural network model with flexible segmentation capabilities across field strengths and resolution; nnU-Net, a DL algorithm trained on pULF reference WML masks; and Pseudo-Label Assisted nnU-Net (PLAn), a DL algorithm pre-trained on HF reference WML masks and refined with 64mT reference WML masks. Two models were trained with nnU-Net, one using T2-FLAIR images only (nnU-Net-FL) and one using T1w and T2-FLAIR images (nnU-Net-FL/T1). The same was done with PLAn, creating PLAn-FL and PLAn-FL/T1. The six automated WML segmentation outputs were compared to the manual segmentations to determine Dice Similarity Coefficient (DSC) scores. Associations of WML volume estimates with clinical measures were investigated. DSC scores with pULF reference WML masks from PLAn-FL (DSC mean {+/-} SD: 0.50 {+/-} 0.24) outperformed MIMoSA (0.24 {+/-} 0.20, p < 0.0001), WMH-SynthSeg (0.30 {+/-} 0.18, p < 0.0001), nnU-Net-FL (0.41 {+/-} 0.24, p < 0.0001), and nnU-Net-FL/T1 (0.41 {+/-} 0.26, p = 0.0004). Worse Expanded Disability Status Scale (EDSS) and Scripps Neurologic Rating Scale (SNRS) scores were correlated with higher WML volumes in the pULF and HF reference masks. They were also correlated with WML volumes derived from WHM-SynthSeg, nnU-Net-FL, nnU-Net-FL/T1, PLAn-FL, and PLAn-FL/T1, but not MIMoSA. After adjusting for age, WHM-SynthSeg, nnU-Net FL, nnU-Net-FL/T1, PLAn-FL, and PLAn-FL/T1 had significant associations with EDSS and SNRS scores. nnU-Net and PLAn performed best in segmenting WML on pULF-MRI at 64 mT, providing accurate quantitative estimates of WML burden. Moreover, WML volumes estimated by these algorithms were associated with clinical measures of disability, underscoring their utility for reflecting clinical and radiological disease severity. Given pULF-MRI's mobility and lower cost, these findings highlight its relevance in clinical trials, particularly in involving more participants who face logistical constraints and barriers.
Pasyar, P.; Mei, K.; Im, J. Y.; Roshkovan, L.; Geagan, M.; Noël, P. B.
Show abstract
ABSTRACT Background: Metallic implants such as orthopedic screws, prostheses, and dental hardware produce beam-hardening, photon-starvation, and streak artifacts that degrade computed tomography (CT) image quality, and the metal artifact reduction (MAR) methods developed to mitigate them require objective, reproducible benchmarking. Purpose: Objective evaluation of MAR algorithms in CT is hindered by the absence of phantoms that simultaneously provide anatomically realistic backgrounds, embedded implants of known geometry, and controllable, ground-truth--referenced artifact intensity. We present a dual-filament, voxel-level three-dimensional (3D) printing method that fulfills these requirements and demonstrate its capabilities on a clinically representative cervical spine case with embedded orthopedic spinal screws. Methods: The proposed method extends the PixelPrint framework, a fused-deposition-modeling (FDM) workflow that converts clinical Digital Imaging and Communications in Medicine (DICOM) data directly into 3D-printer Geometric code (G-code) without intermediate segmentation or surface meshing, to interleaved, voxel-level deposition of two filaments: a calcium-doped polylactic acid (PLA) for soft tissue and bone, and a higher-attenuation metal-doped PLA for metallic implants. For demonstration, anonymized DICOM data of a healthy cervical spine were used to design and fabricate three matched phantoms, each with six embedded spinal screws at C4--C6: a 0% metal-infill ground-truth phantom, a 50% medium-metal-infill phantom, and an 85% high-metal-infill phantom. All phantoms were scanned on a clinical spectral CT system at 120 kVp and 1000 mAs, reconstructed at 0.67 mm slice thickness with virtual monoenergetic imaging (VMI) across 50--190 keV. Method performance was characterized by region of interest (ROI)-based Hounsfield Unit (HU) agreement with the source patient data and by the noise-independent Gumbel-distribution p-index metric. Results: The dual-filament method reproduced patient anatomy, soft-tissue contrast, and screw geometry with high fidelity. ROI HU values agreed with patient data within {+/-}25 HU for soft tissue and trabecular bone; cortical regions were underestimated owing to the current ceiling of the calcium-doped PLA used in this study. The tunable-artifact behavior was quantified as follows: the Gumbel location parameter scaled monotonically from 46.7 HU (no-metal background) to 57.1 HU (50% infill) to 90.5 HU (85% infill) for the VMI 70 keV with standard filter. High-keV VMI reconstructions substantially reduced streak and beam-hardening artifacts while preserving anatomic detail. Conclusions: The proposed dual-filament, voxel-level PixelPrint method enables the fabrication of patient-specific, multi-material CT phantoms with embedded metallic implants and controllable, ground-truth--referenced artifact intensity. Although demonstrated here in a single cervical-spine case, the workflow is anatomy- and implant-agnostic by construction and could in principle be adapted to other musculoskeletal sites (e.g., knee, hip, dental) and implant materials, providing a reproducible methodological foundation for benchmarking MAR algorithms, characterizing spectral CT performance, and validating emerging photon-counting detector systems. Keywords: 3D printing methodology; fused deposition modeling; voxel-level multi-material printing; spectral computed tomography; metal artifact reduction; phantom design; orthopedic implants; dual filament; PixelPrint.
Hsu, C.-Y.; Liu, Q.; Shyr, Y.
Show abstract
As machine learning and artificial intelligence systems are increasingly used in healthcare, rigorous evaluation of their classification performance has become critical. The F1 and F{beta} scores are widely adopted metrics for assessing performance in imbalanced biomedical data. Recently, we introduced psF1, a unified statistical framework for inference and study design for single and comparative F1 and F{beta} scores under the assumption of independent classifiers. In practice, however, benchmarking two classifiers on the same dataset creates a correlated paired setting. Ignoring this intrinsic dependency leads to overestimation of the standard error and a substantial loss of statistical power. To address this, we develop psF1pair, an advanced framework for statistical inference and power analysis that explicitly accounts for correlations between classifier pairs. Extensive simulation studies demonstrate the performance of psF1pair, and its utility is further illustrated through application to a real-world imaging classification system. As expected, higher correlation between classifiers yields narrower confidence intervals and enhanced statistical power. A freely available R package is provided to facilitate implementation, supporting accurate evaluation and study design for predictive and classification models in biomedical research.
Saleh, M. M.; Hegazy, M.; Alsaied, M. A.; Elkenani, A. J.; Ehab, R.; Hesham, M.; Abdelrazek, H. M.; Nazemi, S.; Shalaby, M.; El-Hussuna, A.
Show abstract
Background: KRAS mutation status is an important biomarker in rectal cancer, with implications for prognosis and treatment response. MRI-based radiomics has emerged as a non-invasive approach for predicting tumor genotypes. However, the diagnostic performance of MRI radiomics for predicting KRAS mutation status remains unclear. This study aimed to evaluate the diagnostic accuracy of MRI radiomics for predicting KRAS mutations in rectal cancer. Methods: A systematic search of PubMed, Cochrane Library, Scopus, and Web of Science was performed through July 2025. Diagnostic test accuracy studies evaluating MRI-based radiomics or artificial intelligence models for predicting KRAS mutation status in adult patients with rectal cancer were included, using molecular testing as the reference standard. Risk of bias was assessed using the QUADAS-2 tool. Pooled sensitivity and specificity were estimated using a bivariate random-effects model. Results: Seven studies involving 1,224 patients were included. The pooled sensitivity was 0.736 (95% CI: 0.697-0.772) and the pooled specificity was 0.645 (95% CI: 0.586-0.701). The false positive rate was 0.355 (95% CI: 0.299-0.414). The area under the hierarchical summary receiver operating characteristic curve was 0.754, with a normalized partial AUC of 0.666. Between-study heterogeneity ranged from low to moderate depending on the estimation method (I2 = 8.4%-53.3%). Conclusion: MRI radiomics demonstrates moderate diagnostic accuracy for predicting KRAS mutation status in rectal cancer and may serve as a promising non-invasive biomarker for preoperative molecular stratification. Further large-scale studies with external validation are required to confirm its clinical utility.
Di Giovanni, D. A.; Tanaka, A.; Horikoshi, T.; Tsuboyama, T.; Yokota, H.; Zakarian, R.; Matsumoto, Y.; Vallieres, M.; Reinhold, C.
Show abstract
Purpose: To compare the cross-site generalization of radiomic features and deep learning embeddings for MRI prediction of substantial lymphovascular space invasion (LVSI) in endometrial cancer. Materials and Methods: This retrospective two-center study included 206 women (mean age, 59.8 years) with endometrial cancer who underwent preoperative 3-T MRI from March 2016 to March 2023. Hospital A (n = 130) was used for development and Hospital B (n = 76) for strict external testing. T2-weighted, reduced field-of-view diffusion-weighted, and apparent diffusion coefficient images were manually segmented. Radiomic features and seed-pooled embeddings from 3D ResNet18, DenseNet121, and U-NEXtractor were modeled with elastic-net logistic regression or XGBoost. Out-of-fold Platt calibration and sensitivity-targeted thresholds were estimated using development data only. AUCs were summarized with 95% bootstrap confidence intervals. Results: External radiomics with elastic-net achieved an AUC of 0.609 (95% CI: 0.464, 0.740) and sensitivity of 0 of 12 (0%). DenseNet121 with elastic-net had the highest external AUC (0.685; 95% CI: 0.538, 0.822) but sensitivity of 3 of 12 (25%). U-NEXtractor with elastic-net detected 10 of 12 positive cases (83.3%) with specificity of 32 of 64 (50.0%) and balanced accuracy of 0.667. XGBoost showed higher apparent development performance but weaker external operating behavior. Conclusion: Under real-world cross-site MRI acquisition shift, DenseNet121 and U-NEXtractor embeddings showed better external generalization than handcrafted radiomic features for substantial LVSI prediction.
Amiri, S.; Afshar, P.; Rohban, M. H.
Show abstract
Objectives. Radiomics pipelines extract hundreds of quantitative features that are widely known to be redundant, but the structure of this redundancy is usually treated as a per-dataset nuisance to be pruned away. We tested the alternative hypothesis that a substantial number of feature-feature correlations are universal: they persist across patients and across anatomically distinct structures because they reflect shared mathematical and image-statistical properties of how the image is summarised, rather than properties of the tissue being imaged. Materials and Methods. We re-analysed the publicly available Radiomics Atlas Dataset of normal Abdominal and Pelvic CT (RADAPT), restricting the analysis to the 526 non-contrast-enhanced examinations of the 531-subject atlas and to the 107 original (non-filtered) PyRadiomics features. The 53 segmented structures were grouped into four broad anatomical categories -- bones, muscles, vessels, and parenchymal organs. RADAPT is distributed as one Excel file per structure, with patients as rows and features as columns. Within each structure file we z-score-normalised every feature across patients, computed the absolute Spearman correlation matrix, and retained edges with |{rho}| [≥] {tau} for {tau} in {0.70, 0.80, 0.90}. We then intersected the edge sets across all structure files to obtain a "universal" correlation graph, in which an edge survives only if it exceeds the threshold in every structure (each estimated across the full patient sample). Stable feature communities were defined as the maximal cliques of this graph. Robustness to patient sampling was tested by repeating the entire pipeline on five independent random splits of each file into two patient halves (10 sub-cohorts per threshold), and the implementation was independently reproduced in R. Results. Despite the strictness of the global-intersection criterion, 34, 24, and 14 stable feature communities survived at {tau} = 0.70, 0.80, and 0.90 respectively, with the largest cliques containing six members at {tau} = 0.70 and {tau} = 0.80 and five members at {tau} = 0.90. The community structure was clearly interpretable: separate cliques captured (i) variance-like intensity dispersion, (ii) long-run / low-frequency (coarse) texture, (iii) high gray-level texture, (iv) low gray-level texture, (v) volume and surface shape, and (vi) local-homogeneity and energy/entropy duals. On random-half resampling the exact-match recovery rate of these communities was 81.5 %, 86.7 %, and 80.7 % across the three thresholds; departures from exact recovery were almost always a single boundary feature added or dropped, consistent with finite-sample fluctuation of near-threshold edges rather than structural instability. The R re-implementation reproduced the Python results exactly. Conclusion. A substantial portion of radiomics feature collinearity is universal across patients and tissues. We distinguish two layers within it: trivial near-algebraic duals that are universal by construction, and non-trivial cross-matrix-family communities that are the genuine empirical finding. Together they provide an interpretable, definition-grounded basis for aggressive dimensionality reduction, for retrospectively reconciling apparently different feature selections in the literature, and for moving radiomics pipelines toward organ-agnostic, more reproducible models. Clinical relevance statement. Selecting a single representative feature from each universal community shrinks the original-feature space by roughly an order of magnitude without sacrificing biologically distinct information. For example, the five variance-family members (first-order Variance, GLCM SumSquares, GLCM ClusterTendency, GLDM and GLRLM GrayLevelVariance) can be replaced by a single representative, removing redundant degrees of freedom that would otherwise inflate model variance; and labelling each retained feature by its community lets two studies that selected different variance-family names be recognised as having found the same signal, simplifying model development and improving cross-cohort generalisability in clinical CT workflows.
Rivera, J.; Zhou, Y.; Sak, L.; Pudewa, F.; Lee, J.; Yamamoto, M. T.; Yoo, H.; Lum, M.; Zhang, M.; Patel, A.; Vandenberghe, L. E.; Fenn, S. K.; Wang, Y.; Bailey, B.; Holley, S. M.; Vivas, A. C.; Holly, L. T.; Lu, D. C.
Show abstract
Objective: Photobiomodulation therapy has emerged as a promising modality to facilitate scar healing and pain management in dermatology and plastic surgery. However, its role in postoperative care following spine surgeries remains understudied. This double-blinded, placebo-controlled study aimed to investigate the effects of photobiomodulation in patients with chronic lower back pain undergoing lumbar decompression, with postoperative wound healing as the primary outcome and pain reduction and functional recovery as secondary outcomes. Methods: Patients were randomized to receive either active photobiomodulation braces (N=13) or placebo braces (N=12). Follow-up assessments were performed at 2, 4, 6, 8, and 12 weeks postoperatively. Outcomes included wound healing (Stony Brook Scar Evaluation Scale), back and leg pain (Visual Analog Scale), quality of life (EuroQol 5D), and functional status (Oswestry Disability Index). Results: Compared to the placebo group, the photobiomodulation treatment group had a 4.12-fold cumulative improvement in final scar scores, with significant between-group differences at postoperative weeks 6, 8, and 12 (p = 0.0062, 0.010, 0.042). Among patients with severe preoperative disability, treatment resulted in a 1.89-fold faster improvement in back pain (p=0.025) and a 1.80-fold faster improvement in ODI scores (p=0.025); and superior treatment effect on wound healing were again observed at weeks 6, 8, and 12. Among patients with poor initial scars, treatment led to a significantly better scar outcome than placebo at week 6 and a 1.94-fold faster EQ5D improvement (p=0.052), with significant gains observed as early as two weeks after surgery. There were no adverse events associated with photobiomodulation treatment. Conclusions: Photobiomodulation significantly promoted postoperative wound healing following lumbar decompression surgery, with therapeutic benefits preserved even in patients with poor baseline scar scores and functional impairment. This indicates that the efficacy of photobiomodulation is not limited by the initial scar condition or disability, supporting its broad clinical applicability. Additionally, patients with severe preoperative disability experienced greater benefits from photobiomodulation than placebo, including faster reduction in back pain and more rapid improvement in functional capacity, highlighting its role in postoperative pain management and rehabilitation. These therapeutic effects are likely mediated by photobiomodulation-induced reduction of inflammation and enhancement of tissue repair. Together, this study suggests that photobiomodulation can be a promising adjunct therapy to facilitate postoperative recovery in patients undergoing spine surgery.
Weerasinghe, C.; Osowicki, J.; Simpson, J. A.; Crocker-Buque, T.; McCarthy, J.; Williams, E.; Price, D. J.
Show abstract
Controlled human infection models (CHIMs) are increasingly used in infectious disease research to study pathogen dynamics and evaluate interventions under controlled conditions. However, these studies are resource-intensive and involve ethical and safety constraints, making efficient study design critical. Dose-finding is a key early component in CHIMs, where the aim is to identify a challenge dose that achieves a target infection probability. Traditional rule-based designs are commonly used but can be inefficient, motivating the use of model-based adaptive approaches such as the Bayesian Continual Reassessment Method (CRM). Although CRM has been extensively studied and widely adopted in Phase I oncology trials for identifying the maximum tolerated dose of therapeutics, its application in CHIM settings remains limited, particularly when the endpoint of interest is infection. This tutorial provides step-by-step guidance for implementing a Bayesian CRM in dose-finding CHIMs, using an oropharyngeal Neisseria gonorrhoeae challenge as a motivating case study. The framework outlines key design components, including dose-grid specification, dose-response model, prior elicitation, Bayesian updating, decision rules, and stopping criteria, with particular emphasis on a clinically interpretable parameterisation. Trial operating characteristics are evaluated through simulation studies under multiple dose-response scenarios and prior-predictive analyses, and compared with a commonly used '3+3' type rule-based design. This work highlights the advantages of Bayesian model-based designs for dose-finding in CHIMs over classic rule-based designs and provides a structured, reproducible framework for implementing CRM, supporting their application in future CHIM studies.
Huang, Y.-H.; Arana, K.; Rachimi, S.; Tam, H.; Spegarova, J. S.; Engelhardt, K. R.; Griffin, H.; Mee, M.; Miano, M.; Raggi, F.; Grossi, A.; Rusmini, M.; Ceccherini, I.; Dell'Orso, G.; Ferro, J.; Giarratana, M. C.; Pillai, V.; Banka, S.; Garcez, T.; Briggs, T. A.; Mellouli, F.; von Hardenberg, S.; Beier, R.; Auber, B.; Baumann, U.; Tawamie, H.; Behrens, E.; Oldridge, D. A.; Cabrera, E. C.; Xu, Y.; Ouyang, S.; Hambleton, S.; Romberg, N.; Cyster, J. G.
Show abstract
The X-linked G-protein coupled receptor GPR174 is highly expressed in T and B lymphocytes and has immunoregulatory roles in mice, but its function in humans is unknown. We describe a cohort of six individuals who have function-disrupting variants in GPR174 and a clinical phenotype of lymphadenopathy and autoimmunity. Histological analysis of two patient lymph nodes revealed necrotizing lymphadenitis and lymphoproliferation resembling Kikuchi-Fujimoto disease. In-depth analysis of three patients and related carriers revealed overaccumulation of CD8 terminally differentiated effector memory cells re-expressing CD45RA (TEMRA). Patient cells and GPR174-deficient CD8 T cells generated from controls showed less repression of proliferation by the GPR174 ligand lysophosphatidylserine (lysoPS) and an effector-biased gene expression program. GPR174-deficient CD4 T cells were resistant to lysoPS-mediated suppression of IL2 production. In mice, chronic viral infection led to over-accumulation of GPR174-deficient effector CD8 T cells. We describe an inborn error of immunity associated with dysregulated lymphocyte responses that we propose predisposes to exaggerated lymphoproliferation and autoimmunity following viral infection.
Schiffer, H.; Fisher, L.; Curtis, H. J.; Wood, C.; Brown, A. D.; Bacon, S. C.; Croker, R.; Goldacre, B.; MacKenna, B.; Speed, V.; Macdonald, O.
Show abstract
Lithium has been the gold standard for the treatment and prevention of relapse in bipolar disorder for over 60 years. Guidance from the National Institute for Health and Clinical Excellence states explicitly to 'offer lithium as a first-line, long-term pharmacological treatment for bipolar disorder'. Yet, in the last two decades its use has been in decline with clinicians favouring anticonvulsants or antipsychotics when treating this condition. In this study, we have used three openly available datasets containing prescribing data from primary and secondary care to explore trends in the use of lithium in England, showing both regional and temporal variance between 2015-2024. We have shown that lithium use declined in primary care by 20.9% in the last ten years (2015-2024) and 10.9% overall in the last five years (2019 to 2025). We have also shown how there is some regional variation in the source of lithium for patients, although the vast majority is prescribed in primary care. Further research into clinical behaviour is needed to understand what is driving the decrease in lithium usage, and what barriers and enablers may influence its use across the country.
Arendse, G.; Kamerman, P.; Wadley, A.; Edwards, R. R.; Joska, J.; Parker, R.; Madden, V. J.
Show abstract
Objective: There is a bidirectional relationship between emotional distress and pain. However, this relationship is understudied in people with HIV in low-resource settings. This study sought to describe the temporal relationship between emotional distress and pain in people with HIV. Design: Longitudinal observational study. Methods: Participants with virally suppressed HIV, reporting either no pain or persistent pain at baseline, provided weekly remote ratings of distress, worst pain, and average pain using 0-10 visual analogue scales. Within-individual fluctuations in distress and pain were visualised over time. Group-level correlations were determined using Spearman's correlation tests. Cumulative link mixed models assessed whether distress and pain each predicted the other in the following week. Results: 72 participants provided responses over 49 weeks. The participants had a median (IQR) age of 43 (37-51) years, 63% (n=45) were unemployed and most were females (n=51;71%). Distress and pain fluctuated concurrently within individuals: distress was positively correlated with worst pain ({rho}=0.66, 95% CI= 0.60-0.72, p<0.001) and average pain ({rho}=0.70, 95% CI=0.64-0.75, p<0.001) intensity within the same week. Worst pain (OR=1.42, 95% CI=1.17-1.71, p<0.001) and average pain (OR=1.43, 95% CI=1.20-1.71, p<0.001) intensity both predicted distress in the next week. Distress predicted worst pain intensity (OR=1.25, 95% CI=1.07-1.46, p=0.023) but not average pain intensity (OR=1.19, 95% CI=1.01-1.40, p=0.152) in the next week. Conclusions: The temporal relationship between distress and worst pain intensity was bidirectional, whereas distress did not temporally predict average pain intensity. Both pain and emotional distress should receive attention from HIV research and clinical care in low-resource settings.
Aung, K. W.; Scuffell, J.; Podlasek, A.; Engamba, S.; Jones, F.; Edwards, A.; Chew-Graham, C. A.; Sanyaolu, L.; Busse-Morris, M.
Show abstract
Background Post-infection conditions (PICs), such as Long Covid, are associated with heterogeneous, fluctuating symptoms that profoundly affect daily functioning. Despite moderate-certainty evidence from the NIHR-funded LISTEN trial (COV-LT2-0009) that personalised self management support improves outcomes and may reduce societal and economic impacts of Long Covid, many people living with PICs still receive condition-specific services, generic advice, or stand-alone digital tools that do not address their complex needs. Aim To map care approaches in general practice and synthesise UK evidence for PIC management. Design and setting Scoping review and online survey. Method A two-phase study was conducted: (1) a scoping review of UK evidence on PIC management in general practice; and (2) a supplementary online survey of practitioners working in UK general practice to provide contextual insights. Results The scoping review identified 32 studies focused on Long Covid. One study included a comparator group (ME/CFS). Study populations were predominantly white ethnicity and female. Evidence for non-Covid PICs in UK general practice was largely absent. The supplementary survey (n=46) provided preliminary practice-level insights. Healthcare practitioners reported varied PIC presentations, diagnostic uncertainty, limited referral pathways, inequitable access, and low confidence in managing PICs. Conclusion Evidence informing PIC management in UK general practice remains predominantly Long Covid-focused and may not reflect the range of PICs encountered in practice. While survey findings are preliminary and require confirmation in larger samples, they highlight uncertainty around PIC management. Further research is needed to evaluate whether existing Long Covid pathways should be expanded or complemented by broader PIC models. Keywords general practice; Long Covid; self-management; post-viral syndromes
John, A.; Pike, C.; Olga, L.; Sovio, U.; Wong, H. S.; Smith, G. C.; Aiken, C.
Show abstract
Background: Children born prematurely (before 37 weeks) or admitted to the neonatal unit (NNU) are at increased risk of adverse long-term physical health outcomes. It is also recognised that there is an association with later academic performance and special educational needs, however it is not clear whether these broad risk factors could be used as stand-alone heuristics to identify children who may benefit from additional support in educational settings. We aimed to examine the associations between neonatal unit (NNU) admission and educational attainment in mid-childhood. Methods and Findings: Pregnancy data from a prospective birth cohort (Pregnancy Outcome Prediction Study, Cambridge, United Kingdom, 2008-2012) were linked to national educational outcomes (Department for Education, United Kingdom). Multivariable regression models adjusted for maternal, child, and socioeconomic factors were used to evaluate associations between (i) all NNU admissions, (ii) at term NNU admissions >48 hours, (iii) preterm birth without ongoing physical health needs, and educational outcomes at ages 5-11 years. Children who required any NNU care were more likely not to meet expected educational standards across multiple ages and domains in early and mid-childhood: age 5 early year foundation (aOR 1.64, 95% CI 1.19-2.27, p=0.003), phonics at age 6 (aOR 2.43, 95% CI 1.72-3.57, p<0.001), and at age 7 (here assessments were divided into multiple domains): reading (aOR 1.67, 95% CI 1.18-2.38, p=0.004), writing (aOR 1.72, 95% CI 1.25-2.38, p<0.001), mathematics (aOR 1.56, 95% CI 1.09-2.22, p=0.020), and science (aOR 1.85, 95% CI 1.22-2.78, p=0.003). Similar patterns were observed among both at term-born infants who stayed >48hrs in NNU (phonics assessment at age 6 aOR 2.26, 95% CI 1.51-3.36, p<0.001) and in children born preterm without long-term physical health sequelae (phonics assessment at age 6 aOR 3.07, 95% CI 1.96-4.81, p<0.001). These associations were robust to adjustment for demographic, perinatal, and socio-economic factors. By age 11, differences in academic attainment were attenuated and no longer clearly distinguishable across all exposure groups. However, there was an increased likelihood of special educational needs (SEN) at age 11 associated with any NNU admission (aOR 1.78, 95% CI 1.15-2.73, p=0.009), at term NNU admission for >48hrs (aOR 1.88, 95% CI 1.19-3.00, p=0.007), and children born preterm without long-term physical health sequelae (aOR 1.50, 95% CI 1.00-2.25, p=0.049). Predictive performance of any NNU admission for SEN at age 11 was moderate (AUC 0.70, 95% CI: 1.14-2.65, p=0.010), with balanced sensitivity and specificity and high negative predictive value. Conclusions: NNU admission, for both term and preterm infants, is associated with poorer educational outcomes and an increased likelihood of special educational needs in mid-childhood.
Djaafara, B. A.; Elyazar, I. R.; Yosephine, P.; Surya, A.; Silalahi, F. S.; Handito, A.; Thohir, B.; Aryani, D.; Gunawan, D.; Nisa, A. K.; Prianto, E.; Samad, I.; Cook, A. R.; Huang, A. T.; Clapham, H. E.; Bhatt, S.; Mishra, S.
Show abstract
Estimating dengue force of infection (FOI) is essential for understanding transmission dynamics and targeting intervention programmes, yet surveillance data in endemic settings required for estimations are often incomplete, with varying formats. We developed a Bayesian hierarchical catalytic model that jointly fits age-stratified case data, aggregate case data, and seroprevalence surveys within a single framework, incorporating external covariates to improve parameter identifiability. Synthetic validation showed that covariates alone recovered accurate FOI point estimates even when most districts contributed only aggregate data, but did so with poorly calibrated uncertainty; anchoring the model with a single seroprevalence survey was necessary to bring credible interval coverage close to nominal. Applied to 128 districts across Java and Bali, Indonesia (2016-2024), the model revealed substantial spatial heterogeneity in FOI and reporting rates. Many districts in Java exceeded the WHO-suggested seroprevalence threshold for vaccine introduction, yet were classified as low-priority when using reported incidence as prioritisation criterion, particularly in areas with weak surveillance. Model-based seroprevalence estimation, integrating multiple data sources, offers a more consistent basis for identifying high-priority districts for vaccine introduction, and is less susceptible to surveillance bias than reported incidence.
van Boven, M.; Bootsma, M. C.
Show abstract
Stochastic epidemic models are a cornerstone of infectious disease epidemiology and are often used to study intervention scenarios. However, large run-to-run variability can make intervention effects difficult to estimate precisely. We revisit the epidemic Sellke construction, which assigns each individual an infection threshold for the cumulative infection hazard such that, conditional on the thresholds, the epidemic trajectory becomes deterministic. This enables coupling of simulations with and without an intervention, yielding low-variance effect estimates even when outcomes such as final size or peak incidence vary widely between runs. We develop an exact, event-driven implementation that maintains infection and recovery events in priority queues. Cumulative infection-hazard updates require O(log N) time per event, yielding overall complexity O(Elog N) for E events in a population of size N. The implementation achieves computational performance comparable to the classical Gillespie algorithm while naturally accommodating non-Markovian infectious periods and complex infectiousness profiles. We illustrate the approach using distance-dependent spread of avian influenza between poultry farms in the Netherlands and a multilayer population with households, schools, and workplaces. In both examples, coupling enables efficient within-run comparisons of intervention scenarios across stochastic realisations.
Gu, S.; Petrovitch, D.; Hall, O. T.; Lambert, J. W.; Kember, R. L.; Nahid, N. A.; Ma, Q.; Sprague, J. E.; McDonough, C. W.; Johnson, J. A.
Show abstract
Background: Opioid use disorder (OUD) is heritable, yet most genome-wide association studies (GWAS) have focused on European populations, leaving the genetic architecture of OUD in non-European populations underexplored. Methods: We conducted GWAS of OUD across three ancestries using electronic health records and genomic data from 52,357 All of Us Research Program participants (8,912 cases; 43,445 matched opioid-exposed controls; 48.5% female). Participants were stratified into European (EUR), African (AFR), and Admixed American (AMR) ancestry groups for logistic regression GWAS, with independent replication in the Million Veteran Program. We then applied the deep-learning model AlphaGenome to predict the tissue-specific transcriptomic and splicing consequences of top risk variants across 13 reward-pathway brain regions. Results: We identified and replicated a novel DDX6 risk locus, alongside established OPRM1 and FURIN signals. AlphaGenome predicted the DDX6 regulatory allele downregulates the stress-resistance gene FOXR1 in the nucleus accumbens, while the protective OPRM1 variant (rs1799971) upregulates OPRM1 expression across reward networks. Other signals of interest included IL6R and SHISA9 (EUR); GHR (AFR); and ASTN2 (AMR). Conclusions: This study identifies DDX6 as a novel OUD risk locus, replicates associations with OPRM1 and FURIN, and highlights biologically plausible ancestry-specific signals in AFR and AMR populations. We also replicated top variants in an independent population. Finally, integrating GWAS with deep-learning annotations provides specific, localized biological hypotheses to guide future experimental validation and targeted therapeutics.